Research on Improved Distributed Association Rules Mining Algorithm in Hadoop Cloud Platform

نویسنده

  • Fenghua Liu
چکیده

In this paper, data mining of association rules, data mining of association rules on distributed databases and distributed encrpytion techniques are introduced. Secondly, existing algorithms of data mining of association rules on distributed databases are analyzed in detail, and then they are improved on aspects of efficiency and security, whereafter the algorithm of EP_ DMA is proposed, later some examples are given and then the merit of EP_DMA can be seen. Finally, a frame of system of distributed mining of Association Rules based on EP_DMA is proposed.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Parallel and Scalabale Rules Based Classifier Using Map-reduce Paradigm on Hadoop Cloud

The huge amount of data being generated by today’s data acquisition and processing technologies. Extracting hidden information is become practically impossible from such huge datasets, even then there are several data mining tasks like classification, association rule, clustering, etc. are used for information extractions. Data mining task, classification, consists of identifying a class to a s...

متن کامل

Parallel Rule Mining with Dynamic Data Distribution under Heterogeneous Cluster Environment

Big data mining methods supports knowledge discovery on high scalable, high volume and high velocity data elements. The cloud computing environment provides computational and storage resources for the big data mining process. Hadoop is a widely used parallel and distributed computing platform for big data analysis and manages the homogeneous and heterogeneous computing models. The MapReduce fra...

متن کامل

Research of Fast Search Algorithm Based on Hadoop Cloud Platform

Hadoop is a powerful open cloud computing platform, using MapReduce hardled data seg mentation and merging. Personal data in the flat without any protection, could be attacked at any time, as a result, cloud platform on the personal fast search algorithm problem is very important. In this paper, a Hadoop cloud data search platform to have safety rules was proposed, it increase the safety of the...

متن کامل

Performance Evaluation of Apriori Algorithm on a Hadoop Cluster

Frequent Itemset Mining is a well-known concept in data sciences. If we feed frequent itemset miner algorithms with large datasets they become resource hungry fast as their search space explodes. This problem is even more apparent when we try to use them on Big Data. Recent advances in parallel programming provides good solutions to deal with large datasets but they present their own problems w...

متن کامل

Distributed Sequential Pattern Mining: A Survey and Future Scope

Distributed sequential pattern mining is the data mining method to discover sequential patterns from large sequential database on distributed environment. It is used in many wide applications including web mining, customer shopping record, biomedical analysis, scientific research, etc. A large research has been done on sequential pattern mining on various distributed environments like Grid, Had...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره   شماره 

صفحات  -

تاریخ انتشار 2015